Back

Journal of Clinical and Translational Science

Cambridge University Press (CUP)

Preprints posted in the last 7 days, ranked by how well they match Journal of Clinical and Translational Science's content profile, based on 14 papers previously published here. The average preprint has a 0.02% match score for this journal, so anything above that is already an above-average fit.

1
Increasing Lung Cancer Screening Participation Using an Informational Video Nudge: A Randomized Feasibility Trial

Wain, K. F.; Carroll, N. M.; Maclennan, A. J.; Hixon, B.; Steiner, J.; Ritzwoller, D. P.

2026-09-01 health systems and quality improvement 10.64898/2026.08.28.26361654 medRxiv
Top 0.1%
3.4%
Show abstract

Purpose: Lung cancer screening (LCS) with low-dose computed tomography (LDCT) reduces lung cancer mortality, yet screening participation remains low. We evaluated whether a brief informational video nudge delivered immediately before a scheduled clinical encounter increased LCS ordering and baseline LCS completion. Patients and Methods: We conducted a randomized feasibility trial within Kaiser Permanente Colorado from March through October 2025. LCS-eligible patients with an upcoming primary care or pulmonology appointment were assigned to intervention or usual care based on birth month. Intervention patients were split into two group, a group who received the LCS informational video nudge via text message within 24 hours of an eligible appointment; and second group who received the text plus a QR code video link during appointment rooming. Outcomes included LCS orders, baseline LCS-LDCT completion, and video engagement. Multivariable logistic regression was used to evaluate factors associated with LCS ordering. Results: Among 1,093 patients, 549 were assigned to intervention and 544 to usual care. Intervention patients were more likely to receive an LCS order within 1 day of their appointment (22.6% vs 16.4%; p=.010) and any time during follow-up (32.6% vs 24.1%; p=.002). Baseline LCS-LDCT completion was 51% higher in the intervention group, although the difference was not statistically significant (8.6% vs 5.7%; p=.078). Among the intervention group, 93 individuals (17%) viewed the video, generating 114 total views, and viewers watched an average of 79% of the video. Most views (82.5%) occurred through text-message delivery rather than QR codes. Conclusion: A brief, low-burden LCS informational video delivered immediately before a clinical encounter and integrated into existing workflows significantly increased LCS ordering and was associated with higher screening completion. Timely, scalable digital nudges may provide an effective strategy for improving LCS participation. Based on the observed effectiveness, feasibility, and efficiency of the intervention, KPCO incorporated the behavioral nudge into standard clinical care in February 2026.

2
Integrated MB-PhD training is a long-term investment in the clinician-scientist workforce

Jafree, D. J.; Sun, M.; Stewart, G. W.; Gishen, F.; Swanton, C.; Motallebzadeh, R.; UCL MB-PhD Outcomes Study Group,

2026-08-31 health policy 10.64898/2026.08.26.26361003 medRxiv
Top 0.1%
3.3%
Show abstract

Background: Clinician-scientists translate clinical observation into discovery, trials, and policy, yet this workforce is shrinking across health systems worldwide. Integrated MB-PhD training, pausing medical training to complete a PhD before clinical exposure or specialisation, is one route into this career. We aimed to evaluate the long-term value of MB-PhD training and the barriers to clinical-academic careers these face after graduation. Methods: We evaluated all 131 graduates (29.8% female) who entered the University College London (UCL) MB-PhD programme over a 25-year period (1994-2018). Bibliometric outputs were collated via an inter-linked information system. Concurrently, all 131 graduates were invited to respond to open-ended questions on career benefits and structural barriers; 99 (75.6%) responded, and responses were independently coded into themes, which were then reviewed and confirmed by a Study Group of 107 individuals, including the 91 respondents who agreed to participate further. Results: Graduates produced 5,877 publications (1,141 first-author, 819 corresponding-author), attracting 350,754 citations, with a mean relative citation ratio of 3.30 {+/-} 0.47, approximately three times the field average and sustained across three decades of programme entry. Graduates secured an estimated $157.55 million across 99 grants, released 465 public datasets, and were named investigators on 31 clinical trials across five continents. Among the 99 survey respondents, 49.5% held consultant-grade posts, 72.7% remained research-active, and 25.3% had reached senior academic grade. Open-ended responses were coded into five recurring structural barriers, subsequently confirmed by the Study Group: insufficient protected research time (72.2% of responses), unsupportive training structures and limited career opportunities (36.7%, 24.4% of responses), funding and pay barriers (22.2% of responses), and lack of mentorship or geographical/family constraints (14.4%, 13.3% of responses). Conclusions: Integrated MB-PhD training generates sustained academic productivity and leadership, but structural barriers threaten retention of graduates within clinical-academic careers. Protecting research time, stabilising funding and pay, and reducing geographic instability are needed to retain the clinician-scientists that health systems have already invested in training.

3
A Randomized Controlled Trial Evaluating a Community-Based, Family Network Heart Health Intervention - the SERVE OC Trial: Design, Rationale and Baseline Findings

Boden-Albala, B.; Wing, J.; Landry, M. J.; Castro, M.; Gutierrez, D.; Cardenas, C.; Rousseau, J.; Rahmani, A. M.; Chavez, A.; Ding, X.; Kurzman, A.; Albala, B.

2026-09-02 public and global health 10.64898/2026.08.31.26361871 medRxiv
Top 0.1%
2.6%
Show abstract

Background: Cardiovascular disease (CVD) disproportionately burdens underserved communities, where social determinants of health (SDOH) perpetuate persistent disparities. Family-based interventions leveraging social support represent a promising yet understudied approach. We describe the rationale, design, and methods of the Skills-based Educational strategies for the Reduction of Vascular Events in Orange County (SERVE OC) RCT and present baseline characteristics of enrolled families. Methods: SERVE OC is a 2-arm RCT of 190 Latino and Vietnamese families (486 individuals) randomized to the family-based intervention or individual self-management. The intervention was grounded in social network theory while employing community engaged strategies. Primary outcomes include achieving ideal cardiovascular health (CVH) defined by AHA Life's Essential 8 (LE8) and systolic blood pressure reduction at 12, 24, and 36 months. Baseline assessments include demographics, LE8, psychosocial factors, food security, and SDOH. Descriptive statistics and regression analyses examined cohort characteristics and associations between SDOH, food security, and LE8. Results: Over 83% of participants had suboptimal LE8 scores. Average adult total LE8 scores were 66.61 {plus minus}11.96, with physical activity as the weakest domain, compared to an average of 76.52{plus minus}10.15 in children. Greater SDOH burden and food security were associated with significantly lower odds of ideal CVH and lower LE8 scores respectively. Conclusions: SERVE OC demonstrates the feasibility of enrolling families in community-engaged RCT targeting CVD disparities in underserved population. Baseline findings confirm substantial CVD risk and SDOH burden underscoring the need for multi-level, culturally tailored interventions. Trials results will inform scalable, family-focused strategies for CVD prevention across the life course. Clinical Trial Registration: URL: https://www.clinicaltrials.gov/; Unique Identifier: NCT05641519.

4
AI Video Analysis of Psychomotor Performance in EMS Education: Agreement With Human Evaluators Across Three Skills

Otte, J. H.; Cartagena, A.

2026-08-31 medical education 10.64898/2026.08.26.26361437 medRxiv
Top 0.2%
1.8%
Show abstract

Background. A primary constraint on the capacity of EMS programs to meet industry demand is psychomotor instruction and verification, requiring direct observation of each student by a qualified evaluator. Whether AI video analysis can relieve it is untested; none has been applied to EMS skill examination or compared with human examiners. Objective. To quantify human EMS evaluator inter-rater reliability and evaluate an AI video-analysis platform against it. Methods. In a prospective, fully crossed study, five certified EMS evaluators and an AI platform independently scored identical video-recorded EMT performances of cervical collar application (n=15), bag-valve-mask (BVM) ventilation (n=14), and medical assessment (n=15) on dichotomous checklists with critical-failure criteria. Agreement was assessed at item, score, and decision levels using Fleiss' kappa, Krippendorff's alpha, Gwet's AC1, and ICC(2,1)/ICC(2,k). Results. Human item agreement was moderate (kappa 0.409 to 0.467), as was single-rater reliability (ICC(2,1) 0.539 to 0.694), against good panel reliability (ICC(2,k) 0.854 to 0.919). Recorded pass/fail agreement was fair (kappa 0.297 to 0.388) and critical-failure agreement near zero for two skills (kappa 0.028, 0.119). AI alignment tracked rubric observability rather than task complexity: r = 0.857 (collar, exceeding every human), -0.173 (BVM), 0.664 (medical), and it was most lenient on two skills. Conclusions. Human evaluators are an imperfect standard, especially on critical failures. The AI was a legitimate additional rater where checklist items were discrete and visually verifiable, but not where credit required judging continuous quantities such as ventilation rate, volume, or suction duration. Defensible uses are formative and archival, not summative. These results reflect an early, non-specialist configuration: a baseline, not a limit.

5
A Pragmatic Randomized Trial of an EHR-Integrated Generative AI Chart Summarization Tool for Ambulatory Clinicians

Chin, A. T.; Zhu, N.; Vangala, S.; Woo, H.; Wisk, L. E.; Kingsley, T.; Mafi, J. N.; Lukac, P. J.

2026-08-31 health informatics 10.64898/2026.08.26.26361496 medRxiv
Top 0.3%
1.3%
Show abstract

BACKGROUND Generative AI (genAI) chart summarization tools embedded in electronic health records (EHRs) are being rapidly deployed across U.S. health systems. Although these tools represent a promising solution to alleviate cognitive burdens, their effects have not been examined in randomized-clinical trials (RCTs). METHODS In this pragmatic RCT at a single academic health system, 284 outpatient clinicians across forty-two specialties were assigned 1:1 to Epic's outpatient chart summarization tool or a usual-care control arm over 90 days, from February 23 to May 23, 2026. The primary outcome was physician task load (PTL) adapted for pre-charting. Prespecified exploratory outcomes included additional validated psychometrics as well as usability, safety, and time-based measures. Descriptive statistics included interaction and usage of the tool. RESULTS Of 74,474 AI chart summaries generated, 14.2% were interacted with by a clinician; the proportion of generated summaries interacted with declined from 21.5% in month 1 to 10.5% in month 3, and the proportion of clinicians using the tool at least once per month declined from 88.7% to 66.2%. The adjusted between-arm difference in PTL at follow-up favored the intervention arm (scale 0-400; -27.4; 95% CI, -49.4 to -5.3; P=0.02). Among the Professional Fulfillment Index (PFI; scale 0-4, lower=better) psychometrics, overall burnout (-0.20; 95% CI, -0.38 to -0.01) and work exhaustion (-0.24; 95% CI, -0.47 to -0.02) were lower in the intervention arm, with little difference in overall professional fulfillment (+0.04; 95% CI, -0.16 to 0.25). Charting time per encounter showed no significant between-arm difference during steady state (-1.2 seconds; 95% CI, -19.0 to 16.6). The net promoter score was -22, indicating that on average, clinicians did not recommend the tool. Among free-text respondents, 57.1% reported at least one concern, most commonly tool limitations or inaccurate information. No adverse patient safety events or near-misses were reported. CONCLUSION An EHR-integrated AI chart summarization tool modestly reduced physician task load and was associated with lower burnout, without time savings and against declining engagement. Sustained usage and oversight of reported inaccuracies remain open challenges.

6
Effects of collaborative clinical visit agenda-setting interventions: A systematic review and meta-analysis

Sierpe, A.; Yen, R. W.; Milliman, A.; Cady, E.; Ahn, B.; Dade, A. E.; Devito, A. M.; Eckert, B. A.; Gopalan, V. V.; Krasinski, S. C.; MacMartin, M. A.; Musacchio, S. G.; Zhang, J.; Saunders, C. H.

2026-09-03 medical education 10.64898/2026.08.30.26361729 medRxiv
Top 0.4%
1.1%
Show abstract

Background Agenda-setting is a fundamental patient-centered communication practice in which a clinician works with a patient to elicit, propose, and organize topics for discussion during a clinical encounter. Various agenda-setting interventions have been developed, including patient-facing tools and clinician training, but their effects have not been systematically evaluated. We aimed to determine the effects of these interventions on encounter, patient, care partner, and clinician outcomes. Methods We searched grey literature and seven databases, including PubMed, from inception through July 2025 for randomized and non-randomized comparative studies of interventions designed to promote or improve clinical visit agenda-setting. Two reviewers independently screened articles and extracted data, with a third reviewer resolving conflicts. We assessed risk of bias using RoB 2 for randomized studies and ROBINS-I for non-randomized studies. We conducted random effects meta-analyses when outcomes were sufficiently comparable, assessed heterogeneity using I2, and rated certainty of evidence using GRADE. Post hoc exploratory subgroup analyses examined study design, adjustment status, and intervention structure. Results Twenty-nine articles describing 22 unique studies met the inclusion criteria, including 13 randomized and nine non-randomized studies. Agenda-setting interventions increased the occurrence of agenda-setting (risk ratio 5.43, 95% confidence interval (CI) 2.06 to 14.28, I2=34.6%) and favored the intervention for concerns addressed when measured as a continuous outcome (standardized mean difference (SMD) 0.37, 95% CI 0.16 to 0.57, I2=65.3%) and overall clinician satisfaction (SMD 0.50, 95% CI 0.23 to 0.78, I2=0.0%). There were no clear differences in the number of concerns raised (mean difference (MD) 0.21, 95% CI -0.19 to 0.61, I2=59.6%), visit duration (MD 0.64 minutes, 95% CI -0.83 to 2.12, I2=51.4%), or overall patient satisfaction (SMD 0.05, 95% CI -0.05 to 0.15, I2=47.0%). Potentially important heterogeneity was present for four of these six outcomes. Post hoc exploratory subgroup analyses did not provide clear evidence that effects varied by study design, adjustment status, or intervention structure. Risk of bias was often high, serious, or critical, and certainty of evidence was low or very low for all pooled outcomes. Conclusions To our knowledge, this is the first comprehensive synthesis of clinical visit agenda-setting interventions. Such interventions may increase the occurrence of agenda-setting and the extent to which patient concerns are addressed without increasing visit length. However, the certainty of evidence was low or very low, and the available evidence does not establish a superior intervention structure.

7
Psychosocial Stress and Allostatic Load Among Underrepresented Minority Women with Familial Cancer Risk

Shachar, E. K.; Haas, R.; Rodriguez, V. E.; Lester, J.; Siavoshi, M. A.; Kwan, L.; Niell-Swiller, M.; Spellman, P. T.; Boutros, P. C.; Chang, V. Y.; Karlan, B. Y.

2026-08-31 public and global health 10.64898/2026.08.26.26361226 medRxiv
Top 0.6%
0.6%
Show abstract

Importance: Chronic stress may contribute to adverse health outcomes through cumulative physiologic dysregulation. Allostatic load (AL), a composite measure of multisystem physiologic burden, may capture biologic effects of structural, social, and psychosocial stress not reflected by self-reported measures. Objective: To evaluate racial and ethnic differences in AL among women with familial cancer risk and examine how socioeconomic status, psychosocial factors, clinical characteristics, and health behaviors contribute to variations in AL. Design: Cross-sectional study of underrepresented minority participants enrolled in the HERSTORY cohort from October 2023 through September 2025, with comparison participants from the UCLA ATLAS biobank. Setting: UCLA academic health system. Participants: The study included 303 racially and ethnically diverse female HERSTORY participants aged [≥]35 years with a family history of cancer and matched non-Hispanic White female ATLAS participants (n=709). Exposures: Race and ethnicity, age, neighborhood deprivation, cancer history and stage, depression, perceived stress, cancer worry, and physical activity. Main Outcomes and Measures: The primary outcome was AL, calculated from cardiometabolic and organ-function measures. A secondary index incorporated race- and ethnicity-specific neutrophil-to-lymphocyte ratio (NLR) derived from 326,826 women in the UCLA Health population. Multivariable regression models evaluated factors associated with elevated AL. Results: Compared with matched non-Hispanic White participants, Black and Asian/Pacific Islander HERSTORY participants had significantly higher AL after adjustment. Hispanic/Latina participants did not have significantly elevated AL. Older age, greater area-level socioeconomic deprivation, and depression were independently associated with higher AL. Prior cancer diagnosis, cancer worry and perceived stress were not significantly associated with AL, whereas regular physical activity was associated with lower AL. Among cancer patients, advanced stage was associated with greater AL. Conclusions and Relevance: This study demonstrates elevated AL among understudied racial/ethnic minority groups with familial cancer risk and identifies associations with neighborhood deprivation, depression, and physical activity. The association between cancer stage and AL suggests that physiologic stress may reflect variation in cancer burden. The lack of association with perceived stress and cancer worry further indicates that physiologic and self-reported psychosocial measures capture distinct dimensions of stress. The development of race/ethnicity-specific NLR thresholds derived from large population samples provide a benchmark for future studies.

8
From Housing to Hotspots: Integrating a Housing-Based Measure of Individual Socioeconomic Status with Geospatial Analysis to Target Colorectal Cancer Screening in Rural Communities

Yao, R.; Wi, C.-I.; Beenken, M. J.; Watson, D.; Wheeler, P. H.; Finch, M.; Kelleher, D. P.; Anil, G.; Anderson, T.; Madden, K.; Okuno, S. H.; Odedina, F. T.; Westfall, E. C.; Park, E. Y.; Sharma, P.; Dugani, S.; Foss, R. M.; Hidaka, B. H.; Sosso, J. L.; Sabarish, S.; Singh, G.; Lugo-Fagundo, N.; Howick, J.; Kim, W. R.; Calvin, A. D.; Walker-Mcgill, C. L.; Rennert, L.; Juhn, Y. J.; Cerhan, J. R.; Lynch, B. A.

2026-09-02 public and global health 10.64898/2026.08.28.26361444 medRxiv
Top 0.8%
0.5%
Show abstract

Purpose: This study assesses the association between colorectal cancer (CRC) screening and a validated, housing-based measure of individual-level socioeconomic status (SES, called HOUSES hereafter) within rural communities and determines whether HOUSES-integrated geospatial analysis can be used to tailor interventions. Methods: We used CRC screening data from a subset of Mayo Clinic Midwest patients living in cities without ready access to routine care in the Mayo Clinic Health System in 2019 to represent rural communities. At the individual level, we assessed the association between CRC screening rates and the HOUSES index, adjusting for age, sex, race/ethnicity, comorbidity, distance from home address to clinic, and area deprivation index, using a multilevel mixed-effects logistic regression model. Additionally, we conducted geospatial analysis to examine the correlation between hotspots of 1) lower CRC screening rates and 2) lower SES of the subject population (HOUSES quartile 1). Findings: Among 34,489 individuals (median age 64.0 years, 52.4% female), those with the lowest SES (HOUSES Q1) had 37% lower odds of being CRC screening adherent than those with the highest SES (HOUSES Q4) (adj. OR [95% CI]: 0.63 [0.58-0.69]). In the 14 identified HOUSES Q1 hotspots, there was a significant correlation in counts of HOUSES Q1 and low CRC screening (correlation coefficient=0.81). Conclusion: Lower SES was significantly associated with lower CRC screening among rural populations. HOUSES-enabled geospatial analysis identified geographic hotspots with lower CRC screening rates for targeted interventions to address disparities in CRC screening in rural communities. HOUSES may be a useful digital tool for cancer preventive care and research.

9
Projected Population-Level Impact of Digital Return of Results for Cardiovascular-Kidney-Metabolic Screening at US Blood Donation Centers: A Monte Carlo Simulation Study

Qian, Z.; Khera, A.; Makhnoon, S.; Chapman, B. E.; Bryant, B.; Sayers, M.; Compton, F.; Eason, S.; Xing, C.; Ahmad, Z.

2026-09-03 public and global health 10.64898/2026.09.01.26360806 medRxiv
Top 1%
0.3%
Show abstract

Background. Cardiovascular-kidney-metabolic (CKM) syndrome affects nearly 90% of US adults, yet most individuals at early, modifiable stages remain unidentified outside clinical care. Blood donation centers offer a scalable, non-clinical venue for CKM screening, but the potential benefit of screening in this context remains unclear. We projected the population-level impact of effective digital return of results (ROR) to inform the design of a pragmatic trial. Methods. We developed a Monte Carlo simulation (100,000 iterations) of the incident major adverse cardiovascular events (MACE), end-stage renal disease (ESRD), and type 2 diabetes (T2DM) preventable by ROR-prompted, guideline-concordant follow-up among donors in CKM Stages 1-2. The estimand counts only events averted by donors who act because of ROR; the intervention effect was modeled directly on strictly positive support, and action was translated into prevented events through a hazard-based cumulative-incidence difference that counts each donor at most once. We evaluated 18 design cells (donor volumes 300,000, 1 million, and 8 million/year; 5- and 10-year horizons; action-rate gains of +10, +20, and +30 percentage points [pp]) and, in a complementary two-arm simulation, the assurance (expected power) of detecting the effect in a single deployment. Results. Under the primary +20 pp scenario, ROR at a single large blood center (300,000 donors/year) is projected to prevent a median of 2,201 events (95% uncertainty interval [UI], 1,099-4,364) over 10 years, scaling to 58,526 (29,154-116,769) at the national donor pool. All 18 design cells had strictly positive 95% lower bounds. The number needed to screen was 136 and the screening cost $2,045 per event prevented (at $15/donor), both invariant to donor volume. Impact scaled linearly with volume and effect size but sub-linearly with the horizon. Detection of the effect was effectively certain at gains of +20 pp or larger (assurance [≥]99.6% in every cell and >99.9% in all but the smallest 5-year cell). Conclusions. Even under the conservative scenario, digital CKM ROR at blood donation centers is projected to prevent hundreds to tens of thousands of incident cardiometabolic events at a screening cost per event well within accepted prevention benchmarks, providing prospective, quantitative justification for a pragmatic, randomized evaluation of digital ROR in non-clinical screening settings.

10
Small dried fish as an affordable source of key micronutrients in Madagascar: nutritional benefits and contamination risks

Todimazava, L. D.; Darias, M. J.; Mouquet-Rivier, C.; Mahafina, J.; Lamy, T.

2026-09-01 nutrition 10.64898/2026.08.28.26361595 medRxiv
Top 1%
0.3%
Show abstract

Micronutrient deficiencies are prevalent in Madagascar, where diets rely heavily on starchy staples and access to animal-source foods is limited. Small dried fish (SDF) are widely available, yet their nutritional value and health risks remain poorly documented. We combined market surveys, taxonomic identification, and micronutrient and heavy metal analyses of nine SDF types collected along National Road 7. The samples encompassed 33 fish families, were dominated by small pelagic species (Clupeidae and Engraulidae), and were appreciated by consumers. A daily portion (5 g for infants; 10 g for young children and women of childbearing age) contributed substantially to Recommended Nutrient Intakes (RNIs). Across samples and groups, SDF were rich (>30% of RNI) in selenium and, for infants and young children, in calcium. All samples were a source of (>15% of RNI), or rich in, phosphorus, whereas iron contributions were more variable but often substantial. Several samples exceeded 100% of RNIs for selenium, calcium, iron, or manganese in infants and young children, and some were also sources of magnesium and, less frequently, zinc. Vitamin A was absent from sun-dried samples but detected in a smoked freshwater type. Heavy metal concentrations varied markedly, and portions of several types led to estimated exposures to inorganic arsenic or cadmium exceeding reference values, whereas freshwater species and some pelagic types showed a more favorable nutrition-risk balance. Overall, SDF are affordable, nutrient-dense foods with strong potential to alleviate micronutrient deficiencies in Madagascar, while highlighting the need for type-specific guidance to balance nutritional benefits and contamination risks.

11
Population-based urinary pesticide biomonitoring in rural Wisconsin: Longitudinal patterns and determinants of glyphosate, AMPA, and 2,4-D, and other modern use pesticides

Schultz, A. A.; Lange, M.; Shelton, B.; Meinholz, E.; Esselman, D.; Paulsen, E.; Haban, A.; Kesner, V.; Rowe, M.; Burke, R.; Tisler, C.; Tomasallo, C.

2026-08-31 public and global health 10.64898/2026.08.26.26361410 medRxiv
Top 1%
0.3%
Show abstract

Background: Population-based biomonitoring of contemporary-use pesticides remains limited in the United States, particularly in rural agricultural regions, and few studies have repeated measurements within the same individuals over time. Methods: We analyzed 28 urinary pesticide-related biomarkers among 600 adults from the population-based Survey of the Health of Wisconsin with archived urine collected during 2008-2016; 296 participants provided repeat urine and updated exposure information in 2025. Detection frequencies, co-detection, and within-person detection patterns were characterized. Generalized estimating equations were used for stacked, repeated-measures analyses of factors associated with detection of aminomethylphosphonic acid (AMPA), glyphosate, 2,4-dichlorophenoxyacetic acid (2,4-D), and any of these three. Prospective-only analyses evaluated more detailed agricultural and recent exposure measures. Results: Glyphosate, AMPA, and 2,4-D were detected in 7.7%, 6.2%, and 4.3% of retrospective specimens and 5.4%, 3.1%, and 4.1% of prospective specimens, respectively. Co-detection and persistent detection across the 9 to 17-year interval was rare. In repeated-measures models, greater fruit and vegetable intake, older age, and male sex were associated with higher 2,4-D detection. Lower household income was associated with lower AMPA detection, while afternoon/evening collection was associated with higher AMPA detection. In prospective analyses, working on field-crop agricultural land showed the strongest agricultural associations, particularly for 2,4-D and detection of any of the three pesticides. Associations were not seen with self-reported conventional versus organic produce consumption. Conclusions: Urinary pesticide detections were generally infrequent in this Wisconsin population. Diet and direct agricultural activities may be more informative exposure pathways than residing near cropland or private well drinking-water characteristics.

12
Global Adoption of openEHR Clinical Data Repositories: A Vendor and Community Survey

Kohler, S.; Meyer-Eschenbach, F.; Michelena, X.; Marschollek, M.; Eils, R.

2026-08-31 health informatics 10.64898/2026.08.27.26361529 medRxiv
Top 1%
0.3%
Show abstract

The openEHR standard provides an open, vendor-neutral architecture for clinical data repositories (CDRs), yet its real-world deployment has not been systematically documented. We conducted a dual-perspective survey combining a vendor survey of openEHR CDR providers with a community survey of openEHR practitioners. Eleven vendor organisations reported deployments across 22 countries and over 100 institutions and health regions. A complementary community survey (n=29, 17 countries) provided context on regulatory environments, adoption drivers, and barriers. Combined, the surveys cover 28 countries, 26 of them with a reported openEHR CDR deployment. Three findings emerge: openEHR has achieved national-scale presence through two distinct channels. Through vendor-market convergence, openEHR-based systems cover the majority of regional health authorities without a national mandate, including 19 of 21 Swedish regions, 3 of 4 Norwegian health regions, and 16 of 21 Finnish wellbeing services counties. Through national health record adoption, governments have built or procured national systems on openEHR as their technical foundation, including Ireland, Malta, Greece, Jamaica and Slovenia. Across Europe, this constitutes an openEHR-based interoperability infrastructure already in place across multiple EU member states. We identified no country in which openEHR is named in binding national regulation, creating structural fragility and an unrealised opportunity for alignment with the European Health Data Space (EHDS). Second, 61% of deployments serve primary use only, and 12% support both primary and secondary use. Third, lack of openEHR-specific knowledge is the most consistent adoption barrier across all geographies and deployment tiers. Adoption is driven by practitioner need and innovation, not by regulatory mandate.

13
Are Frontier Large Language Models Safer Than Government-Backed Symptom Checkers for Clinical Self-Triage? A Standardised Vignette Evaluation

Chowdhury, A. R.; Chowdhury, B.

2026-09-02 health informatics 10.64898/2026.09.01.26361908 medRxiv
Top 1%
0.3%
Show abstract

Background: Consumer use of AI chatbots for health advice is rising, yet triage safety relative to established services remains unclear. Australia's Healthdirect, a government-backed symptom checker with 2.4 million uses in FY2024-25, remains unevaluated against frontier large language models (LLMs), and whether premium subscriptions improve triage safety remains unexplored. This study compared the triage accuracy and safety of Healthdirect against six LLM configurations across ChatGPT, Claude, and Gemini, assessed whether paid subscriptions improve triage safety, and characterised each system's error patterns. Methods: Forty-five clinical vignettes from the Semigran et al. benchmark spanning emergency, non-emergent, and self-care categories (15 each) were evaluated across seven systems. Healthdirect was tested following a seven-rule interaction protocol. LLMs were evaluated using first-person patient-language prompts under free-tier and paid-tier conditions. Outcomes were triage accuracy, emergency sensitivity, under-triage, and critical misses, analysed using Cochran's Q, Bonferroni-corrected McNemar tests, Cohen's kappa, and Wilson intervals. Findings: Triage accuracy differed significantly (Cochran's Q = 36.79, p < 0.001). Healthdirect achieved 48.9% accuracy (95% CI 35.0% to 63.0%; kappa = 0.233) versus 73.3% to 86.7% for LLMs (kappa = 0.600 to 0.800). Healthdirect operated under conservative interactive defaults while LLMs received complete information in a single prompt, which may have disadvantaged Healthdirect. Emergency sensitivity was 46.7% versus 80.0% to 86.7% for LLMs. Healthdirect produced two critical misses; no LLM produced any across 270 evaluations (95% CI 0% to 1.4%). When LLMs undertriaged, they recommended GP care rather than self-care. No tier differences were significant (all p > 0.05), and most systems over-triaged self-care cases. Interpretation: Frontier LLMs demonstrated higher triage accuracy and safer error profiles than Healthdirect. All LLMs avoided critical misses; Healthdirect did not. Premium subscriptions did not significantly improve triage safety. These findings support clinical governance decisions about whether LLMs warrant formal evaluation alongside government-backed symptom checkers.

14
Understanding urgent blood-donor mobilisability: a cross-sectional online survey of digitally reachable adults in Ghana

Shen, H.; Agorinya, I. A.; Ayanore, M. A.; Brede, M.; Chapman, A.; Head, M.

2026-08-31 public and global health 10.64898/2026.08.27.26361538 medRxiv
Top 2%
0.2%
Show abstract

Introduction Safe and timely blood availability remains a major global health challenge, especially in low- and middle-income countries. Digital tools may accelerate donor contact, but digital reachability alone does not ensure that people will notice, trust and act on urgent requests to support blood donation efforts. We examined factors associated with anticipated engagement in digitally coordinated urgent blood-donor mobilisation among digitally reachable adults in Ghana. Methods We conducted a cross-sectional online survey from September 2025 to January 2026 across Ghana's 16 regions. Participants were recruited via Facebook advertising and snowball sampling. Factors associated with urgent blood-donor mobilisability were assessed under four criteria: high future-donation willingness; high willingness to install a trusted donation app; high willingness to respond to a trusted urgent-request; and high practical flexibility to leave current activities. Descriptive analyses and multivariable logistic regression examined prevalence and associated factors. Results Among 1,067 participants, 577 (54.1%) met all four criteria. Future-donation willingness (91.8%), trusted-app installation willingness (83.2%) and trusted-request response willingness (82.7%) were common, whereas practical flexibility was lower (66.6%). In the adjusted model, high formal health-system trust (adjusted OR (AOR) 3.95, 95% CI 2.08-7.50), high digital-response readiness (AOR 2.26, 1.66-3.08), previous donation (AOR 1.47, 1.08-2.01), high donation knowledge (AOR 1.42, 1.03-1.97) and willingness to donate to strangers were positively associated with high mobilisability. Women (AOR 0.60, 0.43-0.83), participants reporting a work-schedule barrier (AOR 0.43, 0.29-0.66) and those travelling over 30 min to the nearest healthcare facility at night (AOR 0.66, 0.45-0.96) had lower adjusted odds. Conclusions Digital reachability and stated donation willingness may overestimate the population pool available for emergency donation. Digital blood-donor solutions should consider verifiable health-system requests, account for response readiness and current availability, and connect willing individuals with accessible collection options and transport support where needed.

15
Default-filled outcome labels in a deployed cognitive-screening programme: an operator-level audit and the construction of twenty-four language-model arms

Ji, J.; Sun, Z.; Ying, X.; Hao, J.; Fu, Z.; Shi, D.; Kong, X.; Xu, Y.; Zhang, X.; Du, X.; Zhang, Z.; Liu, X.; Lin, P.; Wang, H.

2026-09-02 health informatics 10.64898/2026.08.28.26361585 medRxiv
Top 2%
0.2%
Show abstract

Background. Routine service databases are attractive sources of training labels for clinical prediction models, but the processes that write those labels are rarely audited before the labels are used. In a deployed community cognitive-screening programme, we audited the routine cognitive-status label, built a matrix of twenty-four model arms over the same patients under a specialist reference standard, and measured what each supervision choice bought or cost. Methods. The study cohort is the 672 individuals whose cognitive status was recorded by a titled (attending-or-above) physician, that record being the reference standard; after holding out one institution entirely, a development panel of 642 individuals at 38 institutions. The routine cognitive-status label these individuals also carry was first audited at the operator level: for each data-entry account we counted diagnoses entered and the proportion recording any impairment, and tested a competing bulk-timestamp explanation. Twenty-four arms span the supervision choices such a programme faces: an incumbent 21-variable logistic regression; local language models (Qwen2.5-1.5B/3B, Qwen3-4B/8B) zero-shot, with chain-of-thought, fine-tuned on physician labels, on routine labels with and without decontamination, or on a proxy scale-band task; preference-optimised (DPO) and reinforcement-trained (GRPO) variants; a proprietary frontier model queried zero-shot; and knowledge distillation of that frontier model into the regression and into the local 4B, using 943 teacher-labelled records from the programme's unlabelled pool. All arms are scored out-of-fold under one five-fold split grouped on registry-resolved institution clusters (no cluster spans a fold); paired contrasts use a 2,000-draw cluster bootstrap. Results. 181 operator accounts (each entering at least 100 diagnoses with zero recorded impairments) account for 45,315 rows - 40.5% of the outcome column; recorded impairment falls monotonically with account volume (15.7% for 1-9 rows to 0.7% for 500-999); a bulk-timestamp explanation was tested and refuted, identifying the write-time column as a migration artefact. Under the specialist standard, no locally fine-tuned arm beat the incumbent regression (AUROC 0.926): physician-label SFT reached 0.924 (4B), DPO 0.881, and GRPO 0.789; the pre-registered two-stage proxy-then-RL recipe was worse than its single-stage contaminated baseline (-0.030, 95% CI -0.077 to -0.004). Chain-of-thought reduced discrimination at every size (-0.072, -0.080, -0.041 at 1.5B/3B/4B; -0.012, n.s., at 8B). The frontier model scored 0.932 (vs. regression +0.007, n.s.). The distilled 4B reached 0.940 - above the incumbent (+0.014, 0.004 to 0.031) and above its own teacher (+0.008, 0.001 to 0.017) - with near-teacher calibration; it reached the teacher's level by 50 teacher labels and changed little beyond 200. Conclusions. The audit and the arm matrix support one deployment recipe: audit the routine label at the operator level before training on it; do not expect fine-tuning, preference optimisation, or reinforcement learning on a few hundred specialist cases to beat a well-calibrated regression; and if a frontier model is available but undeployable, spend a bounded number of queries on it as a labelling instrument and distil. A companion paper uses these frozen predictions to quantify how evaluation design choices compare with model choice.

16
Acute Cardiovascular and Electrocardiographic Effects of Nicotine Pouches: A study protocle for a Randomized, Double-Blind, Placebo-Controlled Crossover Trial (NICOTUNE STUDY)

Khodi Babaroudi, E.; Pham, M. H. X.; Lenz, I. T.; melgaard, e. l. r.; Grand, J.; Hove, J. D.; Seven, E.

2026-08-31 public and global health 10.64898/2026.08.29.26361726 medRxiv
Top 2%
0.1%
Show abstract

Introduction: Nicotine Pouches are increasingly used as a smokeless alternative to cigarettes and other nicotine products, yet their acute cardiovascular effects remain poorly documented. While nicotine's impact on heart rate and electrocardiogram (ECG) parameters is well-documented in smoking, no trials have evaluated these effects specifically for nicotine pouches. Methods: This study is a single-center, double-blind, placebo-controlled, crossover trial which will include 20 healthy adult nicotine users. Participants will undergo three sessions, receiving either a placebo, 6 mg, or 14 mg nicotine pouch in random order. Heart rate obtained by an ECG and various other ECG parameters, vital signs, and subjective symptoms will be measured at baseline, and multiple time points over 30 minutes. Conclusions: This study aims to determine whether nicotine pouches cause acute changes in heart rate, ECG parameters, vital signs, and self-reported symptoms. We hypothesize that higher nicotine pouch does will lead to measurable increases in heart rate and other autonomic effects compared to placebo.

17
The Psychological Footprint of Unruptured Intracranial Aneurysm Discovery

Renedo, D.; Chen, H.; Sheth, K. N.; Gandhi, D.; Malhotra, A.; Matouk, C. C.

2026-08-31 neurology 10.64898/2026.08.25.26361377 medRxiv
Top 2%
0.1%
Show abstract

Background: Unruptured intracranial aneurysms (UIAs) are increasingly identified incidentally, and management balances rupture risk against treatment risk. UIA diagnosis has been linked to psychological distress, but psychotropic medication initiation after UIA discovery has not been compared across the full UIA management spectrum. Methods: We conducted a retrospective cohort study using IBM MarketScan claims (CCAE, MDCD, and MDCR; 2009-2023) among adults with a UIA diagnosis, continuous enrollment for 365 days before and after the index date, and no SAH/rupture on or before the index date. We compared the prevalence of 6 mental-health diagnoses before versus after UIA discovery and used adjusted logistic regression to examine psychotropic medication initiation within 365 days by management strategy (untreated observation as the reference). Results: Among 54,945 patients (untreated, 78.5%; endovascular, 11.3%; clipping, 3.0%; other/uncertain, 7.2%), prevalence of every mental-health diagnosis was higher after UIA discovery, most for depression (+4.6 percentage points) and anxiety (+4.5 points). Medication initiation was most common for benzodiazepines (8.7%). Endovascular treatment was associated with higher adjusted odds of benzodiazepine (aOR, 1.21), SSRI (aOR, 1.20), and sedative-hypnotic (aOR, 1.25) initiation.Surgical clipping demonstrated the broadest association, with higher odds across 5 of 6 classes, including benzodiazepines (aOR, 1.71) and sedative-hypnotics (aOR, 1.86). Benzodiazepines had the lowest 1-year persistence (10.5%) despite being the most commonly initiated class. Findings were consistent across sensitivity analyses, with the exception of the increase in panic disorder, which was no longer observed after applying a 30-day post-index lag. Conclusions: Mental-health diagnoses and psychotropic medication initiation increased after UIA discovery, and medication initiation was most pronounced among patients treated with surgical clipping. These findings support psychological assessment as part of aneurysm management regardless of strategy.

18
Patient safety culture, teamwork, and observed perioperative safety compliance in a high-volume surgical unit

Sahputri, V.; Angeline, A.; Tenggono, E.

2026-09-02 health systems and quality improvement 10.64898/2026.08.31.26361796 medRxiv
Top 2%
0.1%
Show abstract

Perioperative safety checklists standardize critical actions, but reliable completion depends on the surrounding work system and team behavior. We conducted a prospective observational analytic study from April to May 2026 in the central surgical unit of a high-volume public teaching referral hospital in Indonesia to examine whether patient safety culture and teamwork were associated with directly observed perioperative safety compliance and whether teamwork mediated the culture-compliance relationship. Patient safety culture was measured with the Hospital Survey on Patient Safety Culture 2.0, teamwork with a 35-item TeamSTEPPS Teamwork Perceptions Questionnaire research adaptation, and compliance by direct role-based observation using a 45-item checklist derived from the AORN Comprehensive Surgical Checklist. Eighty of 92 recruited professionals contributed 240 person-operation observations across 50 operations. Overall compliance was 74.75%, with sign-out lowest at 70.68%. Patient safety culture was associated with teamwork ({beta} = 0.590; 95% CI 0.510-0.770) and directly with compliance ({beta} = 0.407; 95% CI 0.187-0.712). The teamwork-compliance coefficient was positive ({beta} = 0.285; p = 0.046), but the prespecified percentile 95% CI included zero (-0.045 to 0.517). The indirect effect through teamwork was not supported ({beta} = 0.168; p = 0.079). These findings support a system-level interpretation of perioperative safety and identify learning-oriented responses to error, situation monitoring, and sign-out fidelity as measurable targets for future improvement efforts.

19
Preferences for receiving study results among pregnant women participating in a phase III clinical trial in Papua New Guinea.

Mengi, A.; Bagita-Vangana, M.; Tesine, P.; Laman, M.; Bolnga, J. W.; Ome-Kaius, M.; Kulimbao, J.; Mase, J.; Mal, L. S.; Mnjala, H.; Lee, G.; Cassidy-Seyoum, S. A.; Thriemer, K.; Unger, H. W.

2026-08-31 medical ethics 10.64898/2026.08.27.26361571 medRxiv
Top 2%
0.1%
Show abstract

Disseminating study results to participants is an ethical responsibility for researchers but remains uncommon in low- and middle-income countries, and participants preferences for receiving study results are poorly understood. This study examined study result dissemination preferences among pregnant women in a phase III malaria prevention trial in Papua New Guinea (PNG). Participants completed an interviewer-administered questionnaire (survey) assessing their interest in and motivation for receiving trial results and preferences for dissemination methods and content. Associations between participants characteristics and dissemination preferences were explored using multivariable logistic regression analysis. Of 1172 trial participants, 96.0% (1125/1172) completed the survey, and of these 99.6% (1121/1125) wanted to learn about the trial results. The main motivation factors driving participants interest were an acknowledgment of their contribution to research (51.7%; n=579) and a better understanding of the study (45.0%; n=505). Most participants (78.9%; n=884) wanted to learn about the trial findings through written summary and a group meeting with other participants at the nearest clinic (31.1%, n=349). Multivariable regression analysis indicated that participants from rural/peri-urban clinics were more likely to choose non-electronic media dissemination approaches such as a group meeting as compared to urban-dwelling participants. Frequently selected items (>50% of participants) for content included information regarding good results of the study, purpose of the study, medical treatment advances, results specific to me, and how study was conducted. There was heterogenicity in the preference for dissemination content: compared to urban clinics rural clinics are less likely to want to learn about how and why study was conducted and medical and scientific advances. Overall, the majority wanted to learn about trial results, highlighting the importance of integrating dissemination into research activities in PNG. Variation in preferences for mode and content of dissemination between study clinics suggests that dissemination activities could be tailored to local context and preferences.

20
GLP-1/GIP Uptake, Indication, and Access Pathways Among US Adults in the Understanding America Study

Chaturvedi, R. R.; Gracner, T.; Perez-Arce, F.; Suen, S.-c.; Jin, J.; Orriens, B.; Pacula, R. L.; Sexton Ward, A.; Haile, R.; Kapteyn, A.

2026-09-02 endocrinology 10.64898/2026.08.28.26361368 medRxiv
Top 2%
0.1%
Show abstract

Importance: Evidence on GLP-1/GIP therapies is largely derived from trials enrolling selected populations or medical records that miss utilization outside healthcare channels. No nationally representative cohort has characterized real-world uptake, indications, and access. Objective: To characterize GLP-1/GIP prevalence, indication, clinical profile, and access. Design: Prospective cohort study with three GLP-1/GIP surveillance waves (March 2024, December 2024, October 2025). Setting: The Understanding America Study, an address-based, nationally representative panel of approximately 15,000 US adults aged 18+ years initiated in 2014. Participants: UAS participants responding to at least one surveillance wave (n=9150). Exposures: GLP-1/GIP use status (never vs any use, comprising current and former use), self-reported primary indication (diabetes, weight loss, or other), and access pathway (traditional vs non-traditional). Main Outcomes and Measures: Survey-weighted prevalence of GLP-1/GIP use, overall and by indication and access pathway; sociodemographic, cardiometabolic, treatment, and access characteristics; and smartwatch-derived resting heart rate, heart rate variability, maximum activity heart rate, step count, and sleep duration and variability. Results: Among n=9150 adults (1274 with any use; 60.9% female; median age 53 years), weighted prevalence increased 46%, from 8.2% (March 2024) to 12.0% (October 2025) representing 32 million. Weight-loss indications grew, reaching nearly half of use (4.1% to 5.6%); diabetes-indicated use was stable (5.3% to 5.4%). Users carried high cardiometabolic burden (obesity, 68.2%; diabetes, 53.6%) but diverged by indication: diabetes-indicated users were older (median, 59 vs 49 years), whereas weight-loss-indicated users were more often female (69.9% vs 51.3%) and healthier. One in three users (~9 million) had non-traditional access, especially in weight-loss-indicated users, of whom 33% had no conventional prescription; 41% used compounding, online, or foreign pharmacies; and, 43% lacked coverage. Non-traditional users were five times as likely to report an unlisted, likely compounded formulation (19.8% vs 4.1%). All p<0.05. Conclusions and Relevance: Real-world GLP-1/GIP use has grown rapidly and diversified substantially in indication, access, and population profile. One in 3 users obtained treatment through nontraditional channels largely invisible to claims data, raising long-term safety, efficacy, and coverage questions. GLIMMER provides a public, nationally representative longitudinal evidence base for future payer and provider decisions.